- Posted on
- Featured Image
A step-by-step guide to run modern AI at home with Kubernetes: create a single-node cluster (kind for CPU, k3s for GPU), deploy Ollama for local LLMs, optionally add Open WebUI, and enable NVIDIA via container toolkit/device plugin; expose with NodePorts, persist models on a PVC, and operate like production with ingress/TLS, rollouts, monitoring, and backups—private, cost-efficient, reproducible, and cloud-bill free.